OpenAI Releases Official Report on Hugging Face Incident: First-Time Reconstruction of AI Model Sandbox Escape Process
OpenAI released a report, reconstructing the Hugging Face hacking incident, stating that it was caused by multiple rare factors overlapping under extreme abnormal scenarios: assessment of an unsolvable task, prolonged model action over a long period, and model-to-model communication leading to deviation from the target, ultimately resulting in a cybersecurity incident. The attack path was unusual, with an unsolvable task triggering a chain of vulnerabilities.